AITopics | Nayarit

Collaborating Authors

Nayarit

Towards Dog Bark Decoding: Leveraging Human Speech Processing for Automated Bark Classification

Abzaliev, Artem, Espinosa, Humberto Pérez, Mihalcea, Rada

arXiv.org Artificial IntelligenceApr-29-2024

Similar to humans, animals make extensive use of verbal and non-verbal forms of communication, including a large range of audio signals. In this paper, we address dog vocalizations and explore the use of self-supervised speech representation models pre-trained on human speech to address dog bark classification tasks that find parallels in human-centered tasks in speech recognition. We specifically address four tasks: dog recognition, breed identification, gender classification, and context grounding. We show that using speech embedding representations significantly improves over simpler classification baselines. Further, we also find that models pre-trained on large human speech acoustics can provide additional performance boosts on several tasks.

identification, recognition, vocalization, (16 more...)

arXiv.org Artificial Intelligence

2404.18739

Country:

North America > United States > Michigan (0.04)
North America > Mexico > Tlaxcala (0.04)
North America > Mexico > Puebla > Puebla (0.04)
(2 more...)

Genre: Research Report > New Finding (0.46)

Industry: Health & Medicine (0.68)

Technology:

Information Technology > Artificial Intelligence > Natural Language (1.00)
Information Technology > Artificial Intelligence > Speech (0.89)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (0.47)

Add feedback

Analysis of Systems' Performance in Natural Language Processing Competitions

Nava-Muñoz, Sergio, Graff, Mario, Escalante, Hugo Jair

arXiv.org Artificial IntelligenceMar-7-2024

Collaborative competitions have gained popularity in the scientific and technological fields. These competitions involve defining tasks, selecting evaluation scores, and devising result verification methods. In the standard scenario, participants receive a training set and are expected to provide a solution for a held-out dataset kept by organizers. An essential challenge for organizers arises when comparing algorithms' performance, assessing multiple participants, and ranking them. Statistical tools are often used for this purpose; however, traditional statistical methods often fail to capture decisive differences between systems' performance. This manuscript describes an evaluation methodology for statistically analyzing competition results and competition. The methodology is designed to be universally applicable; however, it is illustrated using eight natural language competitions as case studies involving classification and regression problems. The proposed methodology offers several advantages, including off-the-shell comparisons with correction mechanisms and the inclusion of confidence intervals. Furthermore, we introduce metrics that allow organizers to assess the difficulty of competitions. Our analysis shows the potential usefulness of our methodology for effectively evaluating competition results.

competition, multiaztertest, wordup, (16 more...)

arXiv.org Artificial Intelligence

2403.04693

Country:

North America > Mexico > Aguascalientes (0.04)
Europe > Spain > Aragón (0.04)
Asia > Middle East > Republic of Türkiye > Ordu Province > Ordu (0.04)
(7 more...)

Genre: Research Report > Experimental Study (1.00)

Industry: Consumer Products & Services (0.46)

Technology: Information Technology > Artificial Intelligence > Natural Language (1.00)

Add feedback

Ethical Considerations for Machine Translation of Indigenous Languages: Giving a Voice to the Speakers

Mager, Manuel, Mager, Elisabeth, Kann, Katharina, Vu, Ngoc Thang

arXiv.org Artificial IntelligenceMay-30-2023

In recent years machine translation has become very successful for high-resource language pairs. This has also sparked new interest in research on the automatic translation of low-resource languages, including Indigenous languages. However, the latter are deeply related to the ethnic and cultural groups that speak (or used to speak) them. The data collection, modeling and deploying machine translation systems thus result in new ethical questions that must be addressed. Motivated by this, we first survey the existing literature on ethical considerations for the documentation, translation, and general natural language processing for Indigenous languages. Afterward, we conduct and analyze an interview study to shed light on the positions of community leaders, teachers, and language activists regarding ethical concerns for the automatic translation of their languages. Our results show that the inclusion, at different degrees, of native speakers and community members is vital to performing better and more ethical research on Indigenous languages.

artificial intelligence, machine translation, natural language, (18 more...)

arXiv.org Artificial Intelligence

2305.19474

Country:

Oceania > Australia (0.14)
South America > Bolivia (0.04)
South America > Peru (0.04)
(15 more...)

Genre: Research Report > New Finding (1.00)

Industry:

Law (1.00)
Information Technology > Security & Privacy (1.00)

Technology: Information Technology > Artificial Intelligence > Natural Language > Machine Translation (1.00)

Add feedback

Context Generation Improves Open Domain Question Answering

Su, Dan, Patwary, Mostofa, Prabhumoye, Shrimai, Xu, Peng, Prenger, Ryan, Shoeybi, Mohammad, Fung, Pascale, Anandkumar, Anima, Catanzaro, Bryan

arXiv.org Artificial IntelligenceApr-27-2023

Closed-book question answering (QA) requires a model to directly answer an open-domain question without access to any external knowledge. Prior work on closed-book QA either directly finetunes or prompts a pretrained language model (LM) to leverage the stored knowledge. However, they do not fully exploit the parameterized knowledge. To address this issue, we propose a two-stage, closed-book QA framework which employs a coarse-to-fine approach to extract relevant knowledge and answer a question. Our approach first generates a related context for a given question by prompting a pretrained LM. We then prompt the same LM for answer prediction using the generated context and the question. Additionally, to eliminate failure caused by context uncertainty, we marginalize over generated contexts. Experimental results on three QA benchmarks show that our method significantly outperforms previous closed-book QA methods (e.g. exact matching 68.6% vs. 55.3%), and is on par with open-book methods that exploit external knowledge sources (e.g. 68.6% vs. 68.0%). Our method is able to better exploit the stored knowledge in pretrained LMs without adding extra learnable parameters or needing finetuning, and paves the way for hybrid models that integrate pretrained LMs with external knowledge.

large language model, machine learning, question answering, (19 more...)

arXiv.org Artificial Intelligence

2210.06349

Country:

North America > United States > California > Los Angeles County > Los Angeles (0.15)
North America > United States > Minnesota > Hennepin County > Minneapolis (0.14)
Europe > Italy > Calabria > Catanzaro Province > Catanzaro (0.04)
(16 more...)

Genre: Research Report > New Finding (0.67)

Industry:

Media (1.00)
Leisure & Entertainment (1.00)
Government > Regional Government > North America Government > United States Government (1.00)
Education (0.68)

Technology:

Information Technology > Artificial Intelligence > Natural Language > Question Answering (0.85)
Information Technology > Artificial Intelligence > Natural Language > Large Language Model (0.72)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (0.32)

Add feedback

Generative Long-form Question Answering: Relevance, Faithfulness and Succinctness

Su, Dan

arXiv.org Artificial IntelligenceNov-15-2022

In this thesis, we investigated the relevance, faithfulness, and succinctness aspects of Long Form Question Answering (LFQA). LFQA aims to generate an in-depth, paragraph-length answer for a given question, to help bridge the gap between real scenarios and the existing open-domain QA models which can only extract short-span answers. LFQA is quite challenging and under-explored. Few works have been done to build an effective LFQA system. It is even more challenging to generate a good-quality long-form answer relevant to the query and faithful to facts, since a considerable amount of redundant, complementary, or contradictory information will be contained in the retrieved documents. Moreover, no prior work has been investigated to generate succinct answers. We are among the first to research the LFQA task. We pioneered the research direction to improve the answer quality in terms of 1) query-relevance, 2) answer faithfulness, and 3) answer succinctness.

large language model, machine learning, question answering, (21 more...)

arXiv.org Artificial Intelligence

2211.08386

Country:

North America > United States > California > Los Angeles County > Los Angeles (0.14)
Asia > China > Hong Kong (0.04)
North America > United States > New York (0.04)
(20 more...)

Genre:

Personal (0.92)
Research Report > New Finding (0.67)

Industry:

Government > Regional Government > North America Government > United States Government (1.00)
Government > Military (1.00)
Education (1.00)
(8 more...)

Technology:

Information Technology > Artificial Intelligence > Natural Language > Question Answering (1.00)
Information Technology > Artificial Intelligence > Natural Language > Large Language Model (1.00)
Information Technology > Artificial Intelligence > Natural Language > Information Retrieval (1.00)
(2 more...)

Add feedback

ERICA: Improving Entity and Relation Understanding for Pre-trained Language Models via Contrastive Learning

Qin, Yujia, Lin, Yankai, Takanobu, Ryuichi, Liu, Zhiyuan, Li, Peng, Ji, Heng, Huang, Minlie, Sun, Maosong, Zhou, Jie

arXiv.org Artificial IntelligenceDec-29-2020

Pre-trained Language Models (PLMs) have shown strong performance in various downstream Natural Language Processing (NLP) tasks. However, PLMs still cannot well capture the factual knowledge in the text, which is crucial for understanding the whole text, especially for document-level language understanding tasks. To address this issue, we propose a novel contrastive learning framework named ERICA in pre-training phase to obtain a deeper understanding of the entities and their relations in text. Specifically, (1) to better understand entities, we propose an entity discrimination task that distinguishes which tail entity can be inferred by the given head entity and relation. (2) Besides, to better understand relations, we employ a relation discrimination task which distinguishes whether two entity pairs are close or not in relational semantics. Experimental results demonstrate that our proposed ERICA framework achieves consistent improvements on several document-level language understanding tasks, including relation extraction and reading comprehension, especially under low resource setting. Meanwhile, ERICA achieves comparable or better performance on sentence-level tasks. We will release the datasets, source codes and pre-trained language models for further research explorations.

computational linguistic, relation, representation, (16 more...)

arXiv.org Artificial Intelligence

2012.15022

Country:

North America > United States > Minnesota > Hennepin County > Minneapolis (0.14)
North America > Mexico > Sinaloa > Culiacán (0.06)
Asia > Indonesia > Borneo > Kalimantan > East Kalimantan > Samarinda (0.05)
(14 more...)

Genre: Research Report > New Finding (0.48)

Industry: Education (0.49)

Technology:

Information Technology > Artificial Intelligence > Natural Language > Text Processing (0.93)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (0.68)

Add feedback

Low-pitched, rumbling rocks could help predict when earthquakes strike, research says

The Japan TimesOct-24-2017, 13:01:36 GMT

TEPIC, MEXICO – Rocks under increasing pressure before earthquakes strike send out low-pitched rumbling sounds that the human ear cannot detect but could be used to predict when a tremor will strike, scientists said Monday. Researchers recreated powerful earthquake forces in a laboratory and used high-tech algorithms to pick out the acoustic clues amid all the other noise of a pending quake, according to findings published in Geophysical Research Letters, a journal published by the American Geophysical Union. The sounds are emitted typically a week before an earthquake occurs, so deciphering them would allow scientists to pinpoint the timing of a tremor, the research paper said. Scientists currently can calculate the probability of an earthquake in a particular area but not when it will happen, according to the U.S. Geological Survey. "People have said you can't predict earthquakes. We're now saying we believe for the first time we can predict an earthquake in a laboratory," said Colin Humphreys, professor of materials science at Cambridge University and one of the paper's authors.

artificial intelligence, humphreys, machine learning, (8 more...)

The Japan Times

Country:

North America > Mexico > Nayarit > Tepic (0.27)
Europe > United Kingdom > England > Cambridgeshire > Cambridge (0.27)
North America > United States > New Mexico > Los Alamos County > Los Alamos (0.09)
North America > United States > California (0.07)

Genre: Research Report > New Finding (0.40)

Technology: Information Technology > Artificial Intelligence > Machine Learning (0.82)

Add feedback